Papers with language teaching

4 papers
SchAman: Spell-Checking Resources and Benchmark for Endangered Languages from Amazonia (2022.aacl-short)

Copied to clipboard

Challenge: Spell-checking as a generation task requires large amount of data, which is not feasible for endangered languages such as the languages spoken in Peru.
Approach: They propose to use augmented misspelling data to train neural spell-checking models for four endangered languages of Peru: Shipibo-Koniba, Asháninka, Yánesha, yine .
Outcome: The proposed model achieves better scores in most of the errors and languages in the four indigenous languages of Peru: Shipibo-Koniba, Asháninka, Yánesha, yine.
Teacher Perception of Automatically Extracted Grammar Concepts for L2 Language Learning (2023.findings-emnlp)

Copied to clipboard

Challenge: Language teachers need to be accessible and have the necessary resources to create effective content for their students.
Approach: They propose to extract grammar descriptions from a natural text corpus that answer questions about morphosyntax and semantics from lexical corpus.
Outcome: The proposed method is applied to two Indian languages, Kannada and Marathi, which, unlike English, do not have well-developed resources for second language learning.
Concept Tokens: Learning Behavioral Embeddings Through Concept Definitions (2026.findings-acl)

Copied to clipboard

Challenge: Concept Tokens is a lightweight method that adds a special token to a pretrained LLM . we find that negating the hallucination token reduces hallucines and lowers precision .
Approach: They propose a lightweight method that adds a new special token to a pretrained LLM and learns only its embedding from multiple natural language definitions of a target concept.
Outcome: The proposed method can learn only its embedding from multiple definitions of a target concept . the study shows that it can improve hallucinations and recasting in closed-book questions .
Towards Semantic Tagging for Irish (2024.lrec-main)

Copied to clipboard

Challenge: well annotated corpora have been shown to have great value in linguistic and non-linguistic research . minority languages suffer from fewer available language resources than majority languages . a new method for evaluation of semantic annotation is being developed for Irish .
Approach: They propose to build a tool-set for semantic annotation of Irish using semantic tags . they propose to use a lexicon built from a variety of sources to evaluate the tool .
Outcome: a new method for evaluation of semantic annotation has been developed for Irish . the proposed method has 90% lexical coverage and almost 80% accuracy .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations